NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

On Generalization Bounds for Neural Networks with Low Rank Layers

Pinto, Andrea; Rangamani, Akshay; Poggio, Tomaso (December 2024, Center for Brains, Minds and Machines (CBMM))

While previous optimization results have suggested that deep neural networks tend to favour low-rank weight matrices, the implications of this inductive bias on generalization bounds remain underexplored. In this paper, we apply a chain rule for Gaussian complexity (Maurer, 2016a) to analyze how low-rank layers in deep networks can prevent the accumulation of rank and dimensionality factors that typically multiply across layers. This approach yields generalization bounds for rank and spectral norm constrained networks. We compare our results to prior generalization bounds for deep networks, highlighting how deep networks with low-rank layers can achieve better generalization than those with full-rank layers. Additionally, we discuss how this framework provides new perspectives on the generalization capabilities of deep networks exhibiting neural collapse. Keywords: Gaussian complexity, Generalization bounds, Neural collapse, Low rank layers
more » « less
Full Text Available
Mitigating De-Authentication DoS Attacks in 802.11 via eBPF and XDP

https://doi.org/10.1109/NetSoft64993.2025.11080545

Sangiorgi, Alessandro; Pinto, Andrea; Tourani, Reza; Esposito, Flavio (June 2025, IEEE)

Free, publicly-accessible full text available June 23, 2026
Optimizing Model Pruning in Decentralized Learning Networks with DFL-Trim

https://doi.org/10.1109/NetSoft64993.2025.11080578

Pinto, Andrea; Masci, Alessandro; Sacco, Alessio; Marchetto, Guido; Esposito, Flavio (June 2025, IEEE)

Free, publicly-accessible full text available June 23, 2026
The Janus effects of SGD vs GD: high noise and low rank

Xu, Mengjia; Galanti, Tomer; Rosasco, Lorenzo; Rangamani, Akshay; Pinto, Andrea; Poggio, Tomaso (February 2024, Center for Brains, Minds and Machines (CBMM))

It was always obvious that SGD with small minibatch size yields for neural networks much higher asymptotic fluctuations in the updates of the weight matrices than GD. It has also been often reported that SGD in deep RELU networks shows empirically a low-rank bias in the weight matrices. A recent theoretical analysis derived a bound on the rank and linked it to the size of the SGD fluctuations [25]. In this paper, we provide an empirical and theoretical analysis of the convergence of SGD vs GD, first for deep RELU networks and then for the case of linear regression, where sharper estimates can be obtained and which is of independent interest. In the linear case, we prove that the component $$W^\perp$$ of the matrix $$W$$ corresponding to the null space of the data matrix $$X$$ converges to zero for both SGD and GD, provided the regularization term is non-zero. Because of the larger number of updates required to go through all the training data, the convergence rate {\it per epoch} of these components is much faster for SGD than for GD. In practice, SGD has a much stronger bias than GD towards solutions for weight matrices $$W$$ with high fluctuations -- even when the choice of mini batches is deterministic -- and low rank, provided the initialization is from a random matrix. Thus SGD with non-zero regularization, shows the coupled phenomenon of asymptotic noise and a low-rank bias-- unlike GD.
more » « less
Full Text Available
Privacy and Efficiency of Communications in Federated Split Learning

https://doi.org/10.1109/TBDATA.2023.3280405

Zhang, Zongshun; Pinto, Andrea; Turina, Valeria; Esposito, Flavio; Matta, Ibrahim (October 2023, IEEE Transactions on Big Data)

Full Text Available
Hercules: An Emulation-Based Framework for Transport Layer Measurements over 5G Wireless Networks

https://doi.org/10.1145/3615453.3616516

Pinto, Andrea; Ashdown, Andrew; Bin Hassan, Tanzil; Cheng, Hai; Esposito, Flavio; Bonati, Leonardo; D'Oro, Salvatore; Melodia, Tommaso; Restuccia, Francesco (October 2023, ACM)

Full Text Available
Characterizing Location Management Function Performance in 5G Core Networks

https://doi.org/10.1109/NFV-SDN56302.2022.9974927

Pinto, Andrea; Santaromita, Giuseppe; Fiandrino, Claudio; Giustiniano, Domenico; Esposito, Flavio (November 2022, 2022 IEEE Conference on Network Function Virtualization and Software Defined Networks (NFV-SDN))

Full Text Available

Search for: All records